Tutorials, deep dives and product notes — built for developers.
Anthropic's Claude Opus 5.5 matches Fable 5.1 on most work at $4/$20 per million tokens: 66.4% on Terminal-Bench 4.0, 1,846 GDPval-AA Elo, and 40% lower cost per task than Opus 5. Full benchmark table with footnotes, real tester results, the four breaking API changes, and where GPT-6 Astra still wins.
GPT-6 Astra review: every benchmark, what people built, pricing, and the Critical cybersecurity rollout — with sources.